Name | Version | Summary | date |
---|---|---|---|
document-data-extractor | 1.0.4 | Best open-source document to markdown extractor for LLM training data. Convert PDF, Word, PowerPoint, Excel, images, URLs to clean markdown, JSON, HTML locally. Alternative to Unstructured, Docling, Marker, MarkItDown, MinerU, PaddleOCR, Tesseract | 2025-07-29 08:25:56 |
hour | day | week | total |
---|---|---|---|
80 | 2268 | 10354 | 304111 |